|
|
|
| Home | Wayback Machine | Archive-It | Blog | Heritrix |
| Anonymous User (login or join us) |
Internet Archive crawldata from Aaron Swartz Crawl, captured by crawl345.us.archive.org:aaronswartz from Mon Jan 21 11:44:34 PST 2013 to Mon Jan 21 04:13:21 PST 2013.
This item is part of the collection: Away from Keyboard: Aaron H. Swartz
Identifier: AS-20130121114434-crawl345
Contributor: Internet Archive
Crawljob: aaronswartz
Creator: Internet Archive
Date: 2013
Firstfiledate: 20130121114628
Firstfileserial: 02423
Identifier-access: http://www.archive.org/details/AS-20130121114434-crawl345
Lastdate: 20130121041321
Lastfiledate: 20130121121256
Lastfileserial: 02432
Mediatype: web
Numwarcs: 10
Operator: lekash@archive.org
Scandate: 20130121114628
Scanner: crawl345.us.archive.org
Scanningcenter: sanfrancisco
Sizehint: 10423220381
Sponsor: Internet Archive
Publicdate: 2013-01-21 23:43:14
Addeddate: 2013-01-21 23:43:14
Imagecount: 241156
Keywords: crawldata
| Information | Format | Size |
| AS-20130121114434-crawl345_files.xml | Metadata | [file] |
| AS-20130121114434-crawl345_meta.xml | Metadata | 1.4 KB |
| Other Files | Web ARChive GZ | WARC CDX Index | Item CDX Index | Item CDX Meta-Index | Text |
| AS-20130121114434-02423.warc.gz |
994.1 MB
|
1.7 MB
|
|||
| AS-20130121114434-crawl345.cdx.gz |
12.7 MB
|
||||
| AS-20130121114434-crawl345.cdx.idx |
9.3 KB
|
||||
| AS-20130121114746-02424.warc.gz |
956.7 MB
|
739.4 KB
|
|||
| AS-20130121115016-02425.warc.gz |
958.6 MB
|
1.5 MB
|
|||
| AS-20130121115313-02426.warc.gz |
953.7 MB
|
1.4 MB
|
|||
| AS-20130121115553-02427.warc.gz |
954.7 MB
|
1.5 MB
|
|||
| AS-20130121115824-02428.warc.gz |
973.5 MB
|
1.4 MB
|
|||
| AS-20130121120122-02429.warc.gz |
957.1 MB
|
873.3 KB
|
|||
| AS-20130121120346-02430.warc.gz |
953.7 MB
|
1.6 MB
|
|||
| AS-20130121120650-02431.warc.gz |
1.3 GB
|
1.2 MB
|
|||
| AS-20130121121043-02432.warc.gz |
953.9 MB
|
1.5 MB
|
|||
| MANIFEST.txt |
660.0 B
|